Papers with computational research
Achieving Common Ground in Multi-modal Dialogue (2020.acl-tutorials)
Copied to clipboard
| Challenge: | tutorial focuses on three main topic areas: grounding in human-human communication, dialogue systems and multi-modal interactive systems. |
| Approach: | This tutorial examines the use of computational dialogue research to design grounding modules and behaviors in cutting-edge systems. |
| Outcome: | This tutorial examines the results of recent research on grounding in human-human communication . it shows how these results lead to rich and challenging opportunities for doing grounding more flexible and powerful ways . |
Measuring and Modeling Language Change (N19-5)
Copied to clipboard
| Challenge: | This tutorial will help researchers answer questions fundamental to the social sciences and humanities . |
| Approach: | This tutorial is designed to help researchers answer questions in the social sciences and humanities . it synthesizes recent computational techniques for handling and modeling temporal data . |
| Outcome: | The tutorial will synthesize recent techniques for handling and modeling temporal data, such as dynamic word embeddings, and identify useful tools for social scientists and digital humanities scholars. |
Dreaddit: A Reddit Dataset for Stress Analysis in Social Media (D19-62)
Copied to clipboard
| Challenge: | Existing computational studies on stress only focus on domains such as speech or Twitter . a corpus of social media text is used to identify stress . |
| Approach: | They propose a text corpus of lengthy social media data for detecting stress . they use 190K posts from five different categories of Reddit communities . |
| Outcome: | The proposed corpus of social media data can be used to identify stress . it includes 190K posts from five different categories of Reddit communities . |
Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing research on stereotypical biases ignores literature on them and results in resource wastage. |
| Approach: | They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions. |
| Outcome: | The proposed models can inherit and amplify stereotypes under certain conditions. |
Blow the Dog Whistle: A Chinese Dataset for Cant Understanding with Common Sense and World Knowledge (2021.naacl-main)
Copied to clipboard
| Challenge: | Cant is important for understanding advertising, comedies and dogwhistle politics . currently, there are very few resources available for the research of cant . |
| Approach: | They propose a large and diverse dataset for creating and understanding cant from a computational linguistics perspective. |
| Outcome: | The proposed dataset can be used to test word embedding similarity and pretrained language models. |
Do Models of Mental Health Based on Social Media Data Generalize? (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Existing literature on the validity of proxy-based methods for annotating mental health status in social media has raised new concerns regarding their use in clinical applications. |
| Approach: | They explore the generalization ability of machine learning classifiers trained to detect depression in individuals across multiple social media platforms. |
| Outcome: | The proposed methods show that they can be used to train and analyze large datasets and that they are robust to large dataset sizes. |
HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing approaches to automate computational research use fixed sequential agent pipelines with weak global coordination, which limits their robustness and overall performance. |
| Approach: | They propose a hierarchical multi-agent framework for end-to-end paper reproduction that employs supervisory manager agents to coordinate specialised agents across fine-grained stages. |
| Outcome: | The proposed framework improves the paper2code benchmark and significantly reduces hallucination in the evaluation. |
Trends, Limitations and Open Challenges in Automatic Readability Assessment Research (2022.lrec-1)
Copied to clipboard
| Challenge: | Readability assessment is the task of evaluating the reading difficulty of a given piece of text. |
| Approach: | They examine the common approaches used for automatic readability assessment and identify their shortcomings and some challenges for the future. |
| Outcome: | The proposed models are compared with existing models and are based on existing ones. |
A Survey of Cognitive Distortion Detection and Classification in NLP (2025.findings-emnlp)
Copied to clipboard
| Challenge: | despite momentum in natural language processing, the field remains fragmented . inconsistencies in CD taxonomies, task formulations and evaluation practices limit comparability . |
| Approach: | This review provides a comprehensive review of 38 studies spanning two decades . they map how CDs have been implemented in computational research and evaluate the methods applied. |
| Outcome: | The paper presents the first comprehensive review of 38 studies spanning two decades . it summarises common task setups and highlights persistent challenges to support more coherent research. |
NORMSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly (2023.emnlp-main)
Copied to clipboard
| Challenge: | Existing methods to understand acceptable behavior have focused on a single culture and manually built datasets from non-conversational settings. |
| Approach: | They propose a framework to automatically extract culture-specific norms from multi-lingual conversations. |
| Outcome: | The proposed framework extracts culture-specific norms from multi-lingual conversations. |
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus (2025.acl-long)
Copied to clipboard
| Challenge: | a dataset of over 1.1M podcast transcripts is largely comprehensive of all English language podcasts available through public RSS feeds from May and June of 2020. |
| Approach: | They propose to build a large-scale open dataset of podcast transcripts that includes metadata, speaker roles, audio features and speaker turns for a subset of 370K episodes. |
| Outcome: | The proposed dataset is largely comprehensive of all English language podcasts available through public RSS feeds from May and June of 2020. |